Hierarchical Decomposition of Handwritten Manuscripts Layouts

نویسندگان

Vincent Malleron

Véronique Eglin

Hubert Emptoz

Stéphanie Dord-Crouslé

Philippe Régnier

چکیده

In this paper we propose a new approach to improve electronic editions of literary corpus, providing an efficient estimation of manuscripts pages structure. In any handwriting documents analysis process, structure recognition is an important issue. The presence of variable inter-line spaces, of inconstant base-line skews, overlappings and occlusions in unconstrained ancient 19th handwritten documents complicates the structure recognition task. Text line and fragment extraction is based on the connexity labelling of the adjacency graph at different resolution levels, for borders, lines and fragments extraction.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Handwritten Character Recognition in Ancient Manuscripts

The eigenface method is a technique that has been widely used for facial recognition algorithms. The method relies on the use of a one-dimensional Singular Value Decomposition (SVD), but recently it has been argued that a two dimensional SVD would be more effective. In this study, both methods were applied to a handwritten character recognition problem. The two methods yielded similar accuracy ...

متن کامل

Connected Component Based Word Spotting on Persian Handwritten image documents

Word spotting is to make searchable unindexed image documents by locating word/words in a doc-ument image, given a query word. This problem is challenging, mainly due to the large numberof word classes with very small inter-class and substantial intra-class distances. In this paper, asegmentation-based word spotting method is presented for multi-writer Persian handwritten doc-...

متن کامل

Graph Clustering by Hierarchical Singular Value Decomposition with Selectable Range for Number of Clusters Members

Graphs have so many applications in real world problems. When we deal with huge volume of data, analyzing data is difficult or sometimes impossible. In big data problems, clustering data is a useful tool for data analysis. Singular value decomposition(SVD) is one of the best algorithms for clustering graph but we do not have any choice to select the number of clusters and the number of members ...

متن کامل

READ-BAD: A New Dataset and Evaluation Scheme for Baseline Detection in Archival Documents

Text line detection is crucial for any application associated with Automatic Text Recognition or Keyword Spotting. Modern algorithms perform good on well-established datasets since they either comprise clean data or simple/homogeneous page layouts. We have collected and annotated 2036 archival document images from different locations and time periods. The dataset contains varying page layouts a...

متن کامل

Off-line Arabic Handwritten Recognition Using a Novel Hybrid HMM-DNN Model

In order to facilitate the entry of data into the computer and its digitalization, automatic recognition of printed texts and manuscripts is one of the considerable aid to many applications. Research on automatic document recognition started decades ago with the recognition of isolated digits and letters, and today, due to advancements in machine learning methods, efforts are being made to iden...

متن کامل

ذخیره در منابع من

ذخیره در منابع من قبلا به منابع من ذحیره شده

{@ msg_add @}

با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره شماره

صفحات -

تاریخ انتشار 2009

Hierarchical Decomposition of Handwritten Manuscripts Layouts

نویسندگان

چکیده

منابع مشابه

Handwritten Character Recognition in Ancient Manuscripts

Connected Component Based Word Spotting on Persian Handwritten image documents

Graph Clustering by Hierarchical Singular Value Decomposition with Selectable Range for Number of Clusters Members

READ-BAD: A New Dataset and Evaluation Scheme for Baseline Detection in Archival Documents

Off-line Arabic Handwritten Recognition Using a Novel Hybrid HMM-DNN Model

عنوان ژورنال:

اشتراک گذاری